Papers with human-plus-AI task
Towards Conceptualization of “Fair Explanation”: Disparate Impacts of anti-Asian Hate Speech Explanations on Content Moderators (2023.emnlp-main)
Copied to clipboard
| Challenge: | Recent work at the intersection of AI explainability and fairness has focused on how explanations can improve human-plus-AI task performance . |
| Approach: | They propose to characterize what constitutes an explanation that is itself "fair" they use not just accuracy and label time, but psychological impact of explanations on different groups . |
| Outcome: | The proposed method is based on content moderation of potential hate speech and its differential impact on Asian vs. non-Asian proxy moderators across explanation approaches. |